International Journal of Drug Policy
○ Elsevier BV
Preprints posted in the last 7 days, ranked by how well they match International Journal of Drug Policy's content profile, based on 12 papers previously published here. The average preprint has a 0.02% match score for this journal, so anything above that is already an above-average fit.
Reese, T.; Audet, C.; Ancker, J.; Wright, A.; Marcovitz, D.; Kast, K. A.; Bridges, J.; Tindle, H.; Shah, M.; von Horn, A.; Matheny, M. E.
Show abstract
Introduction: Risk of recurrent opioid use during buprenorphine-naloxone (bup-nx) treatment is dynamic and remains elevated after initiation, with vulnerability shaped in part by treatment intensity and gaps between visits, yet routine outpatient care relies on episodic encounters and retrospective data. This mismatch can delay recognition of emerging instability and limit timely treatment adjustments. This paper reports the development and specification of an intervention strategy to address this mismatch. Methods: We used a structured, multi-phase design process to specify and configure a measurement-based care (MBC) strategy for bup-nx treatment (Bup-MBC) in outpatient addiction clinics through three phases: (1) a systematic review of patient-reported outcome measures (PROMs) for substance use treatment; (2) a qualitative needs assessment using the Theoretical Domains Framework and COM-B (Capability, Opportunity, Motivation-Behavior) model to identify gaps in risk monitoring, agency, and trust; and (3) iterative co-design with multidisciplinary clinicians to refine workflow fit and trust-preserving use of data. Patients informed item and feedback content during the needs assessment but did not participate in the co-design cycles. Results: Bup-MBC integrates (1) brief between-visit PROMs (e.g., withdrawal, craving, adherence); (2) immediate non-punitive patient feedback; (3) clinician-facing summaries and non-directive prompts in the electronic health record (EHR); and (4) an opt-in between-visit outreach pathway with predefined safety triggers, all configured within existing EHR and patient portal infrastructure. It targets patient and clinician capability to recognize changes in risk, opportunity for action through structured monitoring and visit preparation, and trust and agency through non-punitive communication, without adding substantial burden. The full measure set, severity bands, and question-to-action map are provided as supplementary material. Key trade-offs included prioritizing single-item measures for feasibility, balancing opt-in outreach with safety overrides, and assuming routine clinician use of summaries. Conclusion: This development study specifies an EHR-integrated MBC strategy for outpatient bup-nx treatment. As single-center design work with co-design limited to clinicians and delivery contingent on portal or text-message access, its outputs are hypotheses about mechanism and fit rather than demonstrated effects. Feasibility studies are needed to evaluate uptake, acceptability, workflow fit, and effects on treatment.
Humphries, C.; Brett, J.; Gruber, F.; James, E.; McKendrick, T. I.; McNairn, K. C.; Miell, A.; O'Brien, R.; Rahman, F.; Schölin, L.; Stewart, M.; Casey, A.
Show abstract
Objective To measure the accuracy of clinical coding, clinician review, and a locally deployed large language model (LLM) in identifying alcohol, drug, and self-harm involvement in emergency department (ED) attendances, and quantify prevalence. Design Two-phase diagnostic accuracy study. In a validation week, the identification strategies were assessed against a conflict-adjudicated reference standard (n=2,256); the LLM was then applied to n=105,096 annual attendances at the same site. Setting UK Type 1 Emergency Department treating patients [≥]16yrs. Main outcome measures Prevalence quantification compared with the reference standard; sensitivity, specificity, and balanced accuracy of each strategy; monthly identification rates and adjusted annual prevalence. Results The reference standard identified 12.1% of attendances as involving alcohol, drugs, or self-harm (coding 6.0%; clinician 10.0%, LLM 15.6%). LLM balanced accuracy matched or outperformed clinician review in all three domains (alcohol 0.942 v 0.930, p=0.635; drug 0.959 v 0.791, p<0.001; self-harm 0.982 v 0.908, p=0.004). Coding recorded 1.07 domains per identified patient against 1.32 in the reference standard. Adjusted annual prevalence corresponded to 12,890 domain involvements per year not identifiable in coded data. Subdomain classification found at least 81.6% of self-harm attendances required medical assessment for injury or overdose before psychiatric review. Conclusions Clinical coding identified fewer than half of presentations involving alcohol, drugs, and self-harm and rarely captured co-occurring domains; under-recording was present across a full year. A locally deployed LLM generated more complete structured data from existing clinical text within NHS infrastructure, at a scale which is not feasible for manual review.
Ghuman, D.; Achar, T.; Gambhirrao, D.
Show abstract
Background Alcohol-associated injury is a leading cause of emergency department (ED) utilization in the United States and a clinically important driver of preventable morbidity across the adult lifespan. Prior surveillance research has characterized how the rate and severity of alcohol-associated injury vary by patient age, but whether the seasonal timing of injury risk is equally predictable across age groups (a question directly relevant to the timing of clinical screening intensification and public health intervention) has not been formally tested. Methods We conducted a retrospective surveillance analysis of 45,876 alcohol-associated ED visits among adults aged 18 years and older, identified from the National Electronic Injury Surveillance System (NEISS), 2019-2025 (weighted national estimate: 2,092,319 visits), using the structured Alcohol_Involved indicator introduced into NEISS case abstraction in 2019. Patients were stratified by sex and five age groups (18-24, 25-34, 35-49, 50-64, and [≥]65 years). Single-harmonic cosinor (Poisson) regression was used to estimate the seasonal peak day of injury risk (acrophase) for each stratum. To assess reliability, we performed leave-one-year-out jackknife resampling (seven iterations per group), case-resampling bootstrap confidence intervals (1,000 iterations), and likelihood-ratio tests of seasonal-phase interactions. Results Peak injury timing differed significantly across age groups (X^2 [8] = 2356.2, p < .0001). Adults aged 25-64 years showed a highly reproducible early-to-mid-July peak, with jackknife estimates shifting [≤]14 days when any single study year was excluded. Adults aged [≥]65 years showed significant seasonal variation annually (all p < .0001, amplitude comparable to younger groups) but a pooled peak estimate that shifted by up to 100 days across jackknife iterations. Sex-stratified analyses revealed that this instability was driven entirely by females aged [≥]65 years (jackknife range: 332 days, peak consistently in late October through early January) rather than males aged [≥]65 (jackknife range: 31 days, peak consistently in early August). Hospital admission rates increased monotonically with age from 9.0% (18-24 years) to 31.8% ([≥]65 years). Conclusions Alcohol-associated injury follows a reproducible, calendar-stable summer seasonal pattern in adults aged 25-64 years. Among adults [≥]65 years, the previously reported temporal instability is concentrated in the female subgroup, whose seasonal injury risk does not converge on a fixed calendar window. These findings suggest that fixed-calendar prevention and screening strategies are well suited to working-age adults and older men, but older women may require a year-round, individually tailored approach. Keywords: Alcohol-related injury; Emergency department; Seasonality; Age factors; Sex differences; Injury surveillance; Cosinor analysis; Older adults
Thiessen, K. A.; Breslin, F. J.; Kerr, K. L.
Show abstract
Adolescent substance use is a major public health concern due to increased risk of future physical and mental health conditions. Fronto-striatal functioning - particularly regarding inhibition and reward processing - may increase vulnerability to high-risk substance use. However, it remains unclear if these neurobiological differences precede substance use or are consequences of it. The ongoing Adolescent Brain Cognitive Development (ABCD) Study follows over 10000 youth, offering an unprecedented opportunity to longitudinally examine substance use patterns throughout development. We utilized family-clustered time-varying Cox proportional hazard models to prospectively examine main and interaction effects of right Inferior Frontal Gyrus (IFG) inhibitory control and bilateral nucleus accumbens (NAc) reward response, alongside early life adversity and peer substance use as predictors of alcohol and cannabis onset in the ABCD Study. We identified a significant crossover interaction such that left NAc activity had a slight positive association with first full alcoholic drink in the context of higher right IFG activity but a negative association in the context of lower right IFG activity. However, peer alcohol and cannabis use emerged as the strongest predictors of outcomes. Alcohol onset was also more common in females, and early life adversity was associated only with cannabis onset. Findings indicate that interactions between inhibition- and reward-related brain regions may impact risk for early substance use onset, but these effects may be modest relative to socioenvironmental factors. Additionally, divergent alcohol and cannabis findings suggest that risk profiles are substance specific. Peer-focused strategies should be considered in preventive efforts.
Ji, J.; Sun, Z.; Ying, X.; Hao, J.; Fu, Z.; Shi, D.; Kong, X.; Xu, Y.; Zhang, X.; Du, X.; Zhang, Z.; Liu, X.; Lin, P.; Wang, H.
Show abstract
Background. Routine service databases are attractive sources of training labels for clinical prediction models, but the processes that write those labels are rarely audited before the labels are used. In a deployed community cognitive-screening programme, we audited the routine cognitive-status label, built a matrix of twenty-four model arms over the same patients under a specialist reference standard, and measured what each supervision choice bought or cost. Methods. The study cohort is the 672 individuals whose cognitive status was recorded by a titled (attending-or-above) physician, that record being the reference standard; after holding out one institution entirely, a development panel of 642 individuals at 38 institutions. The routine cognitive-status label these individuals also carry was first audited at the operator level: for each data-entry account we counted diagnoses entered and the proportion recording any impairment, and tested a competing bulk-timestamp explanation. Twenty-four arms span the supervision choices such a programme faces: an incumbent 21-variable logistic regression; local language models (Qwen2.5-1.5B/3B, Qwen3-4B/8B) zero-shot, with chain-of-thought, fine-tuned on physician labels, on routine labels with and without decontamination, or on a proxy scale-band task; preference-optimised (DPO) and reinforcement-trained (GRPO) variants; a proprietary frontier model queried zero-shot; and knowledge distillation of that frontier model into the regression and into the local 4B, using 943 teacher-labelled records from the programme's unlabelled pool. All arms are scored out-of-fold under one five-fold split grouped on registry-resolved institution clusters (no cluster spans a fold); paired contrasts use a 2,000-draw cluster bootstrap. Results. 181 operator accounts (each entering at least 100 diagnoses with zero recorded impairments) account for 45,315 rows - 40.5% of the outcome column; recorded impairment falls monotonically with account volume (15.7% for 1-9 rows to 0.7% for 500-999); a bulk-timestamp explanation was tested and refuted, identifying the write-time column as a migration artefact. Under the specialist standard, no locally fine-tuned arm beat the incumbent regression (AUROC 0.926): physician-label SFT reached 0.924 (4B), DPO 0.881, and GRPO 0.789; the pre-registered two-stage proxy-then-RL recipe was worse than its single-stage contaminated baseline (-0.030, 95% CI -0.077 to -0.004). Chain-of-thought reduced discrimination at every size (-0.072, -0.080, -0.041 at 1.5B/3B/4B; -0.012, n.s., at 8B). The frontier model scored 0.932 (vs. regression +0.007, n.s.). The distilled 4B reached 0.940 - above the incumbent (+0.014, 0.004 to 0.031) and above its own teacher (+0.008, 0.001 to 0.017) - with near-teacher calibration; it reached the teacher's level by 50 teacher labels and changed little beyond 200. Conclusions. The audit and the arm matrix support one deployment recipe: audit the routine label at the operator level before training on it; do not expect fine-tuning, preference optimisation, or reinforcement learning on a few hundred specialist cases to beat a well-calibrated regression; and if a frontier model is available but undeployable, spend a bounded number of queries on it as a labelling instrument and distil. A companion paper uses these frozen predictions to quantify how evaluation design choices compare with model choice.
Chowdhury, A. R.; Chowdhury, B.
Show abstract
Background: Consumer use of AI chatbots for health advice is rising, yet triage safety relative to established services remains unclear. Australia's Healthdirect, a government-backed symptom checker with 2.4 million uses in FY2024-25, remains unevaluated against frontier large language models (LLMs), and whether premium subscriptions improve triage safety remains unexplored. This study compared the triage accuracy and safety of Healthdirect against six LLM configurations across ChatGPT, Claude, and Gemini, assessed whether paid subscriptions improve triage safety, and characterised each system's error patterns. Methods: Forty-five clinical vignettes from the Semigran et al. benchmark spanning emergency, non-emergent, and self-care categories (15 each) were evaluated across seven systems. Healthdirect was tested following a seven-rule interaction protocol. LLMs were evaluated using first-person patient-language prompts under free-tier and paid-tier conditions. Outcomes were triage accuracy, emergency sensitivity, under-triage, and critical misses, analysed using Cochran's Q, Bonferroni-corrected McNemar tests, Cohen's kappa, and Wilson intervals. Findings: Triage accuracy differed significantly (Cochran's Q = 36.79, p < 0.001). Healthdirect achieved 48.9% accuracy (95% CI 35.0% to 63.0%; kappa = 0.233) versus 73.3% to 86.7% for LLMs (kappa = 0.600 to 0.800). Healthdirect operated under conservative interactive defaults while LLMs received complete information in a single prompt, which may have disadvantaged Healthdirect. Emergency sensitivity was 46.7% versus 80.0% to 86.7% for LLMs. Healthdirect produced two critical misses; no LLM produced any across 270 evaluations (95% CI 0% to 1.4%). When LLMs undertriaged, they recommended GP care rather than self-care. No tier differences were significant (all p > 0.05), and most systems over-triaged self-care cases. Interpretation: Frontier LLMs demonstrated higher triage accuracy and safer error profiles than Healthdirect. All LLMs avoided critical misses; Healthdirect did not. Premium subscriptions did not significantly improve triage safety. These findings support clinical governance decisions about whether LLMs warrant formal evaluation alongside government-backed symptom checkers.
Witham, M.; Evison, F.; Bellass, S.; Cooper, R.; Gallier, S.; Pretorius, S.; Sapey, E.; Suklan, J.; Sayer, A. A.
Show abstract
Study Objective Little is known about where in hospital care for multiple long-term conditions (MLTC) is delivered. We aimed to describe pathways of care (ward transfers) and outcomes for people admitted to hospital for unscheduled care by MLTC status and other key sociodemographic characteristics. Design and setting Analysis of routinely-collected electronic health records from a large acute UK hospital. Participants Adult unscheduled care admissions from 1st July 2018 to 30th June 2019. The presence of two or more of 59 long-term conditions was ascertained using ICD-10 codes from previous hospital discharges. Main outcome measures Markov state transition probabilities were derived for ward moves and compared for MLTC vs no MLTC, age, sex, ethnicity and neighbourhood deprivation. Outcomes (length of stay, death, readmission, move from definitive ward) and time spent in emergency and assessment departments were compared between subgroups. Results A total of 33,252 adults, mean age 56.0 (SD 21.9) years were analysed; 14,834 (42.4%) had MLTC. People with MLTC were more likely to die in hospital (4.2 vs 1.9%, p<0.001), transfer to internal medicine wards or older peoples medicine wards, were less likely to transfer to surgical wards, had longer median length of stay (1.83 vs 0.69 days, p<0.001), stayed longer in acute medical units (15.5 vs 9.6 hours, p<0.001), and were more likely to move from their definitive ward (18.2 vs 16.4%, p=0.002). Conclusion Unscheduled hospital care pathways are complex and differ for people with MLTC, who have worse outcomes and may be less likely to receive optimal care.
Williams, G. H.; Allen, T.
Show abstract
Urban air pollution remains a significant public health concern, contributing to premature deaths and adverse health outcomes. However, there is little causal research evaluating the effectiveness of policies designed to improve air quality. This study assesses the impact of all three stages of London's Ultra Low Emission Zone (ULEZ) on air pollution, via PM2.5 levels, and respiratory health, via prescription records for bronchodilator and respiratory corticosteroid medications. Analyses are at general practice level, using a generalised synthetic control method to estimate causal impacts. Stage 1 was associated with a statistically significant but negligible 0.77% reduction in PM2.5 levels, with no corresponding change in prescribing. Stage 2 produced a paradoxical 2.69% increase in PM2.5, alongside a 4.44% decrease in inhaled corticosteroid quantity but a 12.51% increase in average daily quantity (ADQ) usage, suggesting a worsening of disease severity among existing patients. Stage 3 yielded a 2.69% PM2.5 reduction and a modest 2.18% decrease in bronchodilator ADQ usage. Spillover effects beyond the ULEZ boundary were statistically significant, but negligible. We find overall that the ULEZ had minimal effects on both air quality and respiratory prescribing across all three stages. These findings provide new insights into the effectiveness of ULEZ policies in reducing air pollution and its associated health impacts, suggesting the zone's effects are considerably smaller than previously reported, and that integration with broader policy measures may be necessary to achieve meaningful public health gains.
Jafree, D. J.; Sun, M.; Stewart, G. W.; Gishen, F.; Swanton, C.; Motallebzadeh, R.; UCL MB-PhD Outcomes Study Group,
Show abstract
Background: Clinician-scientists translate clinical observation into discovery, trials, and policy, yet this workforce is shrinking across health systems worldwide. Integrated MB-PhD training, pausing medical training to complete a PhD before clinical exposure or specialisation, is one route into this career. We aimed to evaluate the long-term value of MB-PhD training and the barriers to clinical-academic careers these face after graduation. Methods: We evaluated all 131 graduates (29.8% female) who entered the University College London (UCL) MB-PhD programme over a 25-year period (1994-2018). Bibliometric outputs were collated via an inter-linked information system. Concurrently, all 131 graduates were invited to respond to open-ended questions on career benefits and structural barriers; 99 (75.6%) responded, and responses were independently coded into themes, which were then reviewed and confirmed by a Study Group of 107 individuals, including the 91 respondents who agreed to participate further. Results: Graduates produced 5,877 publications (1,141 first-author, 819 corresponding-author), attracting 350,754 citations, with a mean relative citation ratio of 3.30 {+/-} 0.47, approximately three times the field average and sustained across three decades of programme entry. Graduates secured an estimated $157.55 million across 99 grants, released 465 public datasets, and were named investigators on 31 clinical trials across five continents. Among the 99 survey respondents, 49.5% held consultant-grade posts, 72.7% remained research-active, and 25.3% had reached senior academic grade. Open-ended responses were coded into five recurring structural barriers, subsequently confirmed by the Study Group: insufficient protected research time (72.2% of responses), unsupportive training structures and limited career opportunities (36.7%, 24.4% of responses), funding and pay barriers (22.2% of responses), and lack of mentorship or geographical/family constraints (14.4%, 13.3% of responses). Conclusions: Integrated MB-PhD training generates sustained academic productivity and leadership, but structural barriers threaten retention of graduates within clinical-academic careers. Protecting research time, stabilising funding and pay, and reducing geographic instability are needed to retain the clinician-scientists that health systems have already invested in training.
Schultz, A. A.; Lange, M.; Shelton, B.; Meinholz, E.; Esselman, D.; Paulsen, E.; Haban, A.; Kesner, V.; Rowe, M.; Burke, R.; Tisler, C.; Tomasallo, C.
Show abstract
Background: Population-based biomonitoring of contemporary-use pesticides remains limited in the United States, particularly in rural agricultural regions, and few studies have repeated measurements within the same individuals over time. Methods: We analyzed 28 urinary pesticide-related biomarkers among 600 adults from the population-based Survey of the Health of Wisconsin with archived urine collected during 2008-2016; 296 participants provided repeat urine and updated exposure information in 2025. Detection frequencies, co-detection, and within-person detection patterns were characterized. Generalized estimating equations were used for stacked, repeated-measures analyses of factors associated with detection of aminomethylphosphonic acid (AMPA), glyphosate, 2,4-dichlorophenoxyacetic acid (2,4-D), and any of these three. Prospective-only analyses evaluated more detailed agricultural and recent exposure measures. Results: Glyphosate, AMPA, and 2,4-D were detected in 7.7%, 6.2%, and 4.3% of retrospective specimens and 5.4%, 3.1%, and 4.1% of prospective specimens, respectively. Co-detection and persistent detection across the 9 to 17-year interval was rare. In repeated-measures models, greater fruit and vegetable intake, older age, and male sex were associated with higher 2,4-D detection. Lower household income was associated with lower AMPA detection, while afternoon/evening collection was associated with higher AMPA detection. In prospective analyses, working on field-crop agricultural land showed the strongest agricultural associations, particularly for 2,4-D and detection of any of the three pesticides. Associations were not seen with self-reported conventional versus organic produce consumption. Conclusions: Urinary pesticide detections were generally infrequent in this Wisconsin population. Diet and direct agricultural activities may be more informative exposure pathways than residing near cropland or private well drinking-water characteristics.
Khodi Babaroudi, E.; Pham, M. H. X.; Lenz, I. T.; melgaard, e. l. r.; Grand, J.; Hove, J. D.; Seven, E.
Show abstract
Introduction: Nicotine Pouches are increasingly used as a smokeless alternative to cigarettes and other nicotine products, yet their acute cardiovascular effects remain poorly documented. While nicotine's impact on heart rate and electrocardiogram (ECG) parameters is well-documented in smoking, no trials have evaluated these effects specifically for nicotine pouches. Methods: This study is a single-center, double-blind, placebo-controlled, crossover trial which will include 20 healthy adult nicotine users. Participants will undergo three sessions, receiving either a placebo, 6 mg, or 14 mg nicotine pouch in random order. Heart rate obtained by an ECG and various other ECG parameters, vital signs, and subjective symptoms will be measured at baseline, and multiple time points over 30 minutes. Conclusions: This study aims to determine whether nicotine pouches cause acute changes in heart rate, ECG parameters, vital signs, and self-reported symptoms. We hypothesize that higher nicotine pouch does will lead to measurable increases in heart rate and other autonomic effects compared to placebo.
Bastien, J.; Garcia, K.; Wallace, A. L.; Sullivan, R. M.; Hoh, E.; Wade, N. E.
Show abstract
Background: As cannabis policy changes in the United States, secondhand cannabis smoke (SCS) is increasingly common, including within families. However, prevalence of exposure and clinical correlates over time in adolescents are not fully understood. Objectives: (1) To estimate the prevalence of SCS and personal cannabis use in US-based teens exposed to SCS, and (2) examine the cognitive trajectories of adolescents exposed to SCS compared to non-exposed peers. Methods: Data from the Adolescent Brain Cognitive Development (ABCD) Study was used. Participants (n=11,316 of full cohort with follow-up data; n=776 with self-reported family SCS exposure) attended yearly visits from ages 11-17, completing substance use interviews, toxicological testing, and the NIH Toolbox Cognitive battery. Youth with SCS but no personal cannabis use (n=419; 47% female) were matched on prenatal substance exposure, family substance use history, and sociodemographics to non-SCS exposed and non-cannabis-using youth with a 1:2 ratio (Controls n=838). Linear mixed-effects models assessed cognitive performance by SCS*age interactions, accounting for random effects of subject and family. Covariates included sex and alcohol, nicotine, and other substance use. Secondary models analyzed performance by cumulative waves of reported SCS exposure interacting with age. Results: Of the full cohort, 6.9% (n=776) reported exposure to SCS. Of these individuals, 46% endorsed lifetime personal cannabis use by age 17, relative to 20% of non-SCS exposed youth (OR=3.83[95%CI:3.29,4.44]). Within matched participants, SCS*age demonstrated a significant interaction on attention and inhibitory control ({beta}=-0.32, p=.028), with SCS demonstrating reduced improvement over time. More waves of exposure were also associated with worse performance over time ({beta}=-0.39, p=.057). Discussion: Almost half of those who had been exposed to SCS endorsed personal cannabis use. Cognitive findings were domain specific, similar to findings in secondhand tobacco: SCS exposed youth showed restricted improvement in attention and inhibitory control by age 17. Public health and policymakers should make efforts to curb youth SCS exposure, given the potential for risk which has not been fully explored to date.
Garcia Campos, M. A.; Rocha, T. A. H.; Perez de Souza, J. V.; Murase, L. S.; Murta, F.; Sartim, M. A.; Sachett, J.; Seabra de Farias, A.; Azevedo Machado, V.; Wen, F. H.; Staton, C. A.; Monteiro, W. M.; Gerardo, C. J.; Nickenig Vissoci, J. R.
Show abstract
Background: Snakebite envenoming is a major cause of preventable death and disability in the Brazilian Amazon, where long distances, sparse roads, and dependence on river transport delay access to antivenom. We developed location-allocation models to identify community health centers that could strategically expand access to antivenom in Amazonas State, Brazil. Methodology/Principal Findings: We conducted an ecological geospatial study using a 2025 WorldPop population surface, locations of existing and candidate health facilities, and a multimodal road-and-river transportation network derived from OpenStreetMap and HydroSHEDS. Population demand was represented by 7,065 populated centroids, including 1,586 within Indigenous territories. We applied a maximize-coverage algorithm with a six-hour travel-time threshold. Two models were developed: one for Amazonas excluding Manaus and one for populations living in Indigenous territories. Both models began with 77 facilities already providing antivenom and progressively added candidate community health centers until coverage gains plateaued. The plateau occurred at 110 facilities, corresponding to 33 additional centers. In the model excluding Manaus, this configuration covered 1,118,831 people, or 75.11% of the target population; 87.61% of those covered could reach care within three hours. In Indigenous territories, coverage increased from 50.55% to 69.50%, reaching 50,434 people, of whom 81.39% were within three hours of care. Validation used 3,595 snakebite notifications from the 30 highest-burden municipalities in the Brazilian Notifiable Diseases Information System during 2023-2025. The median proportion reaching care within six hours was 40.81% in observed data and 72.17% in model estimates. Conclusions/Significance: Strategically equipping 33 additional existing community health centers could substantially expand timely access to antivenom, particularly in rural and Indigenous areas. Location-allocation modeling that incorporates river transportation can support evidence-based decentralization of time-sensitive health services in geographically complex settings.
Wojcik, S.; Rulkiewicz, A.; Domienik-Karłowicz, J.
Show abstract
Large language models perform well on medical examinations, but users routinely challenge their answers and invoke professional roles, and it is unclear what a system does when a medical credential and a stated task-specific accuracy point in opposite directions. In a factorial experiment on 480 items from four Polish specialty examination sets and three consumer large language model systems (ChatGPT, Claude, Gemini), each item and system received eleven independent conversations. Conditions crossed attributed source role (medical student, experienced specialist), stated prior accuracy on similar questions (2/10, 8/10) and suggestion correctness. The primary outcome was adoption of a prespecified incorrect option when the baseline answer matched the official key, comparing a specialist described as 2/10 with a student described as 8/10. Baseline agreement with the key was 87.2% across 15,683 analyzable conversations. The incorrect option was adopted more often from the specialist described as 2/10 than from the student described as 8/10 (10.2% vs. 7.6%; adjusted risk difference +2.82 percentage points, 95% CI +0.65 to +4.99). Estimates varied across the three systems and only one system-specific interval excluded zero. In a prespecified exploratory analysis with a shared eligibility rule, correct suggestions were adopted far more often than incorrect ones (risk difference +35.7 percentage points, 95% CI +30.8 to +40.7), indicating selective rather than indiscriminate compliance. An incorrect suggestion from a specialist with low stated accuracy was therefore slightly more influential than the same suggestion from a student with high stated accuracy, although the difference was modest and varied across systems. Agreement reached only after a user has disclosed a preferred answer should not automatically be treated as an independent second opinion, and medical large language model systems should be evaluated on how they revise answers after such disclosure, not solely on initial accuracy.
Khan, Z.; McCarthy, C.; Dalton, K.; Jungo, K. T.; Doherty, A. S.; Reeve, E.; Moriarty, F.
Show abstract
Background: Adverse drug withdrawal events (ADWEs) are a key safety concern during deprescribing but remain poorly explored in pharmacovigilance systems. Objectives: To identify and compare ADWE signals across drug classes, different drugs within drug classes, and across patient characteristics, countries, and over time. Methods: A case/non-case disproportionality analysis was conducted in FDA-FAERS and EMA-EudraVigilance pharmacovigilance databases, with stratification by age (adults: 18-64, older adults: [≥]65), sex (male/female), reporting time (2004-2023 in 5-year intervals), and country (for EMA data). Disproportionality analysis (quantitative signal detection) was used to detect signals between ADWEs and drugs using the proportional reporting rate (PRR[≥]2), reporting odds ratio (ROR>1), and information component (IC>0) with case count [≥]5. Results: Overall, 158,501 reports (FDA-FAERS 145,514; EMA-EudraVigilance 12,987) included drug-event pairs related to ADWEs. In FDA-FAERS, clobetasone (IC=5.58; PRR=79.18; ROR=176.90) showed the strongest ADWE signals, followed by hydromorphone (4.85; 29.94; 37.37), hydrocodone, and paroxetine. In EMA-EudraVigilance, ethyl loflazepate (IC=6.01; PRR=119.80; ROR=197.53), clobetasone (5.39; 102.73; 155.10), veralipride, and levomethadone had the strongest signals. Most drugs maintained positive ADWE signals in analysis stratified into adults and older adults. However, among the top 10 drugs (based on highest IC values), buprenorphine/naloxone, desvenlafaxine, and baclofen in FDA-FAERS (ICs 4.95-6.05) showed stronger signals in older adults. A sex-based difference was observed, with paroxetine, venlafaxine, and buprenorphine/naloxone showing a stronger positive signal in females in both databases, whereas several opioids had stronger signals in males versus females across both databases. Conclusion: This study suggests ADWE signals for some medications differ by age and sex, potentially indicating different risks for withdrawal effects.
Packard, S. E.; Russo, T.; Parrott, J.; Sisti, J.; Lans, A.
Show abstract
Objectives: To estimate the prevalence of Post-Exertional Malaise (PEM) among adults with prior COVID-19 and associated mental health and disability outcomes. Methods: We conducted a cross-sectional analysis of data from a survey of 9,620 adults with prior COVID-19 in New York City, collected May - June 2024. PEM was measured with the DePaul Symptom Questionnaire - Post Exertional Malaise, categorized by symptom duration (< 14 vs. [≥]14 hours). Weighted prevalence estimates were stratified by socio-demographic and clinical characteristics. Modified Poisson regression was used to assess the association of PEM with depression, anxiety, and disability. Results: The prevalence of PEM symptoms was 20.9% overall and 4.0% with symptom duration [≥]14 hours, representing over 800,000 New Yorkers affected and over 150,000 who meet a diagnostic criterion for ME/CFS. PEM prevalence was higher among women, transgender and non-binary adults, people of color, and lower educational attainment, chronic comorbidities, or disabilities. PEM was associated with 3 - 4 times higher prevalence of mental health outcomes and 4 - 5 times higher disability scores. Conclusions: PEM symptoms were common and strongly associated with disability and adverse mental health. Screening, pathways to care, and supportive policies are needed to mitigate long-term consequences, particularly among marginalized populations.
MURHABAZI BASHOMBWA, A.; TCHIO-NIGHIE, K. H.; NANA DJAPOU, M. C.; BUH NKUM, C.; BLAMA ABBA, I.; BEKOLO, C. E.; ATEUDJIEU, J.
Show abstract
Health facilities (HFs) routinely administer medicines and are expected to ensure patient safety by detecting, reporting, investigating, and analysing adverse events following exposure to drugs (AEFED). This study aimed to assess the implementation of pharmacovigilance activities in referral and regional health facilities in Cameroon and to identify pharmacovigilance training needs among healthcare personnel (HP). This was a cross-sectional descriptive study targeting referral and regional health facilities and healthcare personnel involved in patient care and pharmacovigilance activities in Cameroon. Health facilities were selected using stratified purposive sampling, while healthcare personnel were selected through exhaustive sampling. Data were collected using semi-structured electronic questionnaires administered face-to-face by trained enumerators. The questionnaires assessed the organization, resources, and implementation of pharmacovigilance activities at health facilities, as well as healthcare personnel knowledge of pharmacovigilance concepts, previous training, and perceived training needs. Of the 14 eligible health facilities, 10 (71.4%) consented to participate in the study. Of the 10 health facilities, 4 (40.0%) had an established pharmacovigilance unit, while 3 (30.0%) reported conducting neither detection nor notification activities. Among the 261 healthcare personnel approached, 214 (81.9%) participated. Only 41.6% had needed knowledge to detect an adverse event, while 72.9% were aware of adverse event notification procedures. Previous exposure to pharmacovigilance training was reported by 37.9% of healthcare personnel, and all participants expressed a need for additional training, particularly on national pharmacovigilance regulations (69.2%), organization of the pharmacovigilance system (67.3%), and adverse event detection (67.3%). The main reported challenges by healthcare personnel in the implementation of pharmacovigilance activities included insufficient budget allocation, limited access to pharmacovigilance training, lack of pharmacovigilance guidelines and insufficient qualified human resources. Pharmacovigilance implementation in referral and regional health facilities in Cameroon remains limited, with gaps in organizational structures, resources, healthcare personnel knowledge, and training. Strengthening pharmacovigilance systems through improved facility capacity, availability of essential tools, and targeted healthcare personnel training is needed to enhance drug safety surveillance.
Tang, P.; Lu, M. W.-H.; Yeung, K.-T.; Guo, B. J.; Wei, K.-F. N.
Show abstract
Background Global labor migration from LMIC to higher-income destinations has expanded rapidly, placing increasing pressure on destination-country health. Existing research on cross-border migrant workers has focused largely on occupational health, general healthcare utilization, and disease-specific risks, while there is considerably less evidence on their sexual and reproductive health. This study contributes to this understudied field by examining the policy and health-system factors that shape the sexual and reproductive health services for migrant workers in Taiwan. Methods A qualitative study was conducted in Taiwan between November 2025 and August 2026. 22 stakeholders were purposively recruited from academia, healthcare, nongovernmental organizations, government, labor brokerage, and employers. Data were collected through semi-structured interviews and small focus groups. Interviews were conducted in Mandarin Chinese, transcribed verbatim, and translated into English. Data were analyzed using framework analysis combining deductive coding based on the AAAQ framework with inductive coding of implementation and contextual themes. Results Gaps were identified across all four AAAQ dimensions. Participants described limited migrant-responsive SRH programming; physical, financial, administrative, social, and information barriers; shortcomings in linguistic and cultural responsiveness; and weaknesses in interpretation, coordination, and continuity of care, despite generally favorable views of Taiwan's clinical quality. Conclusions Our findings show that broad insurance coverage and strong clinical capacity do not by themselves ensure the realization of migrant workers' SRHR. In Taiwan, rights were mediated through labor brokerage, gendered live-in work arrangements, and fragmented governance across health, labor, immigration, and social-welfare systems. Improving migrant SRHR therefore requires stronger implementation of existing protections, reduced dependence on informal intermediaries, and more integrated institutional responsibility for cross-sector migrant health needs.
Chakraborty, R.; Rosenberg, M.; Weigel, M. M.; Pettifor, A.; Kahn, K.; Gomez-Olive, F. X.
Show abstract
Purpose Despite the high documented prevalence of hunger and poor mental health in adolescent girls and young women (AGYW) in South Africa, this relationship remains understudied, with existing studies limited by their cross-sectional designs. This longitudinal study aimed to identify the association of hunger trajectories with anxiety and depressive symptoms, and hope in AGYW. Methods We used secondary data from the HIV Prevention Trials Network (HPTN) -068 conducted in rural Agincourt, South Africa. Complete data from 1779 AGYW collected at baseline (2011/12) and three annual follow-up visits were used. Hunger trajectories, measured using the Household Hunger Scale, were estimated via Group-Based Trajectory Modelling. Self-reported incident anxiety and depressive symptoms and hope were assessed based on AGYWs last two follow-up visits. Covariate adjusted modified Poisson regression models estimated the association between hunger trajectories and incident anxiety symptoms, incident depressive symptoms, and hope. Results Moderate-severe hunger was prevalent in 11.0%, 10.8%, and 6.0% of the households at baseline, follow-up 1, and 2, respectively. Incident anxiety symptoms were reported by 4.5%, incident depressive symptoms by 20.0% and hopelessness by 52.8% of the AGYW. Two hunger trajectories were identified- no hunger (82%) and marginal hunger (18%). Hunger trajectories were not associated with incident anxiety symptoms [RR:1.09, 95% CI: 0.55, 2.18], incident depressive symptoms [RR: 0.97; 95% CI: 0.72, 1.33] nor hope [RR: 1.00; 95% CI: 0.81, 1.23] in AGYW. Conclusion Better understanding of the factors that promote resiliency and mental health of AGYW in this setting is warranted to inform the design of interventions.
Gaidica, M.; Rosengart, M.
Show abstract
Light reaching the retina is a primary regulator of human circadian physiology, acting largely through melanopsin-expressing retinal ganglion cells with peak short-wavelength sensitivity. Delivering known, repeatable retinal doses outside the laboratory is difficult because conventional light sources leave viewing geometry, gaze, and ambient conditions uncontrolled. Consumer extended-reality (XR) glasses fix a bright binocular display in constant geometry relative to the eye, but their suitability as calibrated photic stimulators has not been established. Here we validate a commercial micro-OLED XR display (VITURE Luma Ultra) for controlled retinal photostimulation. A purpose-built host application renders exact 8-bit RGB stimuli while independently controlling hardware brightness and logging all intensity-determining state; spectral radiance was measured at the retinal position of a 3D-printed phantom head with an open-source miniature spectroradiometer, anchored to absolute units by a luminance transfer calibration. The blue primary peaks at 461 nm (FWHM 43 nm), is spectrally invariant across a >10-fold intensity range, and at maximum output delivers an estimated 299 lx melanopic equivalent daylight illuminance, above consensus daytime recommendations, while remaining roughly two orders of magnitude below photobiological safety limits. The red primary is visually effective with minimal melanopic drive (melanopic DER 0.10), enabling spectrally shifted evening stimulation. Unlike the immersive virtual-reality headsets previously used for calibrated light delivery, the see-through form factor preserves the wearer's view of the surroundings--relevant for clinical monitoring in supervised settings such as the intensive care unit. These results show that consumer XR glasses can serve as a dose-calibrated platform for wearable photostimulation using an open-source measurement chain, and provide groundwork for application-layer dose-response studies.